NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

A survey of computational approaches for characterizing microbial interactions in microbial mats

https://doi.org/10.1186/s13059-025-03634-2

Perillo, Vanesa L; Nute, Michael; Sapoval, Nicolae; Curry, Kristen D; Golia, Logan; Yin, Yongze; Ogilvie, Huw A; Nakhleh, Luay; Segarra, Santiago; Bhaya, Devaki; et al (December 2025, Genome Biology)

Full Text Available
Bayesian Filtering on Graphs

https://doi.org/10.1109/ICASSP49660.2025.10887710

Das, Bishwadeep; Navarro, Madeline; Segarra, Santiago; Isufi, Elvin (April 2025, IEEE)

Full Text Available
Fair CoVariance Neural Networks

https://doi.org/10.1109/ICASSP49660.2025.10888282

Cavallo, Andrea; Navarro, Madeline; Segarra, Santiago; Isufi, Elvin (April 2025, IEEE)

Full Text Available
Low-Rank Tensors for Multi-Dimensional Markov Models

https://doi.org/10.1109/ICASSP49660.2025.10889291

Navarro, Madeline; Rozada, Sergio; Marques, Antonio G; Segarra, Santiago (April 2025, IEEE)

Full Text Available
Online Network Inference from Graph-Stationary Signals with Hidden Nodes

https://doi.org/10.1109/ICASSP49660.2025.10890878

Buciulea, Andrei; Navarro, Madeline; Rey, Samuel; Segarra, Santiago; Marques, Antonio G (April 2025, IEEE)

Full Text Available
Redesigning graph filter-based GNNs to relax the homophily assumption

https://doi.org/10.1109/ICASSP49660.2025.10889874

Rey, Samuel; Navarro, Madeline; Tenorio, Victor M; Segarra, Santiago; Marques, Antonio G (April 2025, IEEE)

Full Text Available
Efficient Path Planning with Soft Homology Constraints

https://doi.org/10.1109/CCTA60707.2024.10666586

Taveras, Carlos A; Segarra, Santiago; Uribe, César A (August 2024, IEEE)

Full Text Available
Graph-based self-supervised learning for repeat detection in metagenomic assembly

https://doi.org/10.1101/gr.279136.124

Azizpour, Ali; Balaji, Advait; Treangen, Todd J; Segarra, Santiago (July 2024, Genome research)

Repetitive DNA (repeats) poses significant challenges for accurate and efficient genome assembly and sequence alignment. This is particularly true for metagenomic data, where genome dynamics such as horizontal gene transfer, gene duplication, and gene loss/gain complicate accurate genome assembly from metagenomic communities. Detecting repeats is a crucial first step in overcoming these challenges. To address this issue, we propose GraSSRep, a novel approach that leverages the assembly graph's structure through graph neural networks (GNNs) within a self-supervised learning framework to classify DNA sequences into repetitive and non-repetitive categories. Specifically, we frame this problem as a node classification task within a metagenomic assembly graph. In a self-supervised fashion, we rely on a high-precision (but low-recall) heuristic to generate pseudo-labels for a small proportion of the nodes. We then use those pseudo-labels to train a GNN embedding and a random forest classifier to propagate the labels to the remaining nodes. In this way, GraSSRep combines sequencing features with predefined and learned graph features to achieve state-of-the-art performance in repeat detection. We evaluate our method using simulated and synthetic metagenomic datasets. The results on the simulated data highlight our GraSSRep's robustness to repeat attributes, demonstrating its effectiveness in handling the complexity of repeated sequences. Additionally, our experiments with synthetic metagenomic datasets reveal that incorporating the graph structure and the GNN enhances our detection performance. Finally, in comparative analyses, GraSSRep outperforms existing repeat detection tools with respect to precision and recall.
more » « less
Full Text Available
SC-MAD: Mixtures of Higher-Order Networks for Data Augmentation

https://doi.org/10.1109/ICASSP48485.2024.10447519

Navarro, Madeline; Segarra, Santiago (April 2024, IEEE)

Full Text Available
GIST: distributed training for large-scale graph convolutional networks

https://doi.org/10.1007/s41468-023-00127-8

Wolfe, Cameron R; Yang, Jingkang; Liao, Fangshuo; Chowdhury, Arindam; Dun, Chen; Bayer, Artun; Segarra, Santiago; Kyrillidis, Anastasios (October 2024, Journal of Applied and Computational Topology)

Abstract The graph convolutional network (GCN) is a go-to solution for machine learning on graphs, but its training is notoriously difficult to scale both in terms of graph size and the number of model parameters. Although some work has explored training on large-scale graphs, we pioneer efficient training of large-scale GCN models with the proposal of a novel, distributed training framework, called . disjointly partitions the parameters of a GCN model into several, smaller sub-GCNs that are trained independently and in parallel. Compatible with all GCN architectures and existing sampling techniques, (i) improves model performance, (ii) scales to training on arbitrarily large graphs, (iii) decreases wall-clock training time, and (iv) enables the training of markedly overparameterized GCN models. Remarkably, with , we train an astonishgly-wide 32–768-dimensional GraphSAGE model, which exceeds the capacity of a single GPU by a factor of$$8\times $$ $8 \times$ , to SOTA performance on the Amazon2M dataset.
more » « less
Full Text Available

« Prev Next »

Search for: All records